Papers by Nam Le Hai

6 papers
On the Impacts of Contexts on Repository-Level Code Generation (2025.findings-naacl)

Copied to clipboard

Challenge: CodeLLMs are widely used for code generation, but their ability to handle repository-level dependencies remains underexplored.
Approach: They propose a benchmark for evaluating repository-level code generation based on dependency contexts.
Outcome: The proposed model improves dependency handling and introduces a new metric, Dependency Invocation Rate (DIR), to measure context utilization.
Improving Vietnamese-English Cross-Lingual Retrieval for Legal and General Domains (2025.naacl-short)

Copied to clipboard

Challenge: Existing document retrieval systems focus on a single language, targeting resource-rich languages like English or Chinese.
Approach: They propose auxiliary loss function and symmetrical training strategy for cross-lingual retrieval between Vietnamese and English . they propose a dataset that covers the general domain and extends to the legal field .
Outcome: The proposed dataset significantly improves state-of-the-art models on cross-lingual retrieval tasks.
Enhancing Discriminative Representation in Similar Relation Clusters for Few-Shot Continual Relation Extraction (2025.naacl-long)

Copied to clipboard

Challenge: Existing methods for relation extraction (RE) fail to address the problem of similar relations, which contributes to catastrophic forgetting.
Approach: They propose a relation extraction method that utilizes relation descriptions and dynamic clustering to identify similar relations.
Outcome: The proposed method mitigates catastrophic forgetting and outperforms state-of-the-art methods by a large margin.
MemORAI: Memory Organization and Retrieval via Adaptive Graph Intelligence for LLM Conversational Agents (2026.findings-acl)

Copied to clipboard

Challenge: Existing graph-based memory systems suffer from information dilution, absent provenance tracking, and uniform retrieval that ignores query context.
Approach: They propose a framework that integrates memory organization and retrieval via a Graph Intelligence framework.
Outcome: Evaluated on LOCOMO and LongMemEval benchmarks, MemORAI achieves state-of-the-art performance in memory retrieval and personalized response generation.
Mitigating Non-Representative Prototypes and Representation Bias in Few-Shot Continual Relation Extraction (2025.acl-long)

Copied to clipboard

Challenge: Existing methods for few-shot continual relation extraction (FCRE) face two main challenges: non-representative prototypes and representation bias.
Approach: They propose to use General Orthogonal Frame to create robust class prototypes . they also utilize label description representations as global class representatives .
Outcome: The proposed method outperforms state-of-the-art methods on well-known benchmarks on well known FCRE benchmarks.
MaGiX: A Multi-Granular Adaptive Graph Intelligence Framework for Enhancing Cross-Lingual RAG (2025.findings-emnlp)

Copied to clipboard

Challenge: Recent advances in Graph-based RAG (GRAG) frameworks focus on knowledge graphs for cross-lingual retrieval.
Approach: They propose a new GRAG framework for cross-lingual question answering . MaGiX constructs a multi-granular cross-linguistic knowledge graph using fine-grained attribute descriptions and cross-synonym edges.
Outcome: The proposed framework outperforms prior GRAG systems in retrieval accuracy and generation quality.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations